Audio-Visual Speech Separation and Dereverberation With a Two-Stage Multimodal Network
نویسندگان
چکیده
منابع مشابه
Speech Separation with Dereverberation-based Pre-processing Incorporating Visual Cues
separation with dereverberation-based pre-processing incorporating visual cues.
متن کاملPerceptual Improvement of a Two-Stage Algorithm for Speech Dereverberation
This paper presents three effective proposals for a twostage algorithm for one-microphone reverberant speech enhancement. The original algorithm is divided into two blocks: one that deals with the coloration effect, due to the early reflections, and the other for reducing the long-term reverberation. The proposed modifications consider changing the linearprediction model order, the adaptation s...
متن کاملAuxiliary Multimodal LSTM for Audio-visual Speech Recognition and Lipreading
The Aduio-visual Speech Recognition (AVSR) which employs both the video and audio information to do Automatic Speech Recognition (ASR) is one of the application of multimodal leaning making ASR system more robust and accuracy. The traditional models usually treated AVSR as inference or projection but strict prior limits its ability. As the revival of deep learning, Deep Neural Networks (DNN) be...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
ژورنال
عنوان ژورنال: IEEE Journal of Selected Topics in Signal Processing
سال: 2020
ISSN: 1932-4553,1941-0484
DOI: 10.1109/jstsp.2020.2987209